Skip to content
  • 0 Votes
    1 Posts
    3 Views
    A
    <p>Ayon sa pinakabagong panuntunan sa billing ng OpenAI, gumagamit ang GPT-5.6 Sol at iba pang modelo ng tiered pricing para sa mga request na may napakahabang context.</p><p>Kapag lumampas sa 272K tokens ang context ng isang request, maaaring singilin ang input, cached reads, at output ayon sa mas mataas na tier ng presyo. Nagmumula ang panuntunang ito sa opisyal na pricing ng OpenAI at hindi ito pansamantalang dagdag-singil o maling pagsingil ng AI-ROUTER.</p><p></p><p>Halimbawa, gamit ang standard pricing ng GPT-5.6 Sol:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Item Sa loob ng 272K Lampas sa 272K</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ━━━━━━━━━━ ━━━━━━━━━━━━━━━━━━━ ━━━━━━━━━━━━━━━━━</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Input $5 / 1M tokens $10 / 1M tokens</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ────────── ─────────────────── ─────────────────</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Cached reads $0.50 / 1M tokens $1 / 1M tokens</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ────────── ─────────────────── ─────────────────</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Output $30 / 1M tokens $45 / 1M tokens</code></pre><p></p><p>Ang aktuwal na gastos ay kakalkulahin pa rin batay sa partikular na grupo, account multiplier, at iba pang configuration sa billing.</p><p></p><p>Ang paglitaw ng x2 marker sa usage record ay nangangahulugang na-trigger ng request na ito ang tiered pricing para sa long context. Pangunahing ipinapahiwatig ng marker na pumasok sa mas mataas na tier ang input o cached reads; hindi ito nangangahulugang basta na lang nadoble ang lahat ng gastos sa buong request. Maaaring gumamit ang output ng ibang tier multiplier.</p><p></p><p><strong> ## Paano maiiwasang ma-trigger ang tiered pricing</strong></p><p>Kung hindi kailangan ang 1M-token context window, inirerekomendang ibalik ang context limit ng Codex sa loob ng 272K.</p><p></p><p>Buksan ang:</p><p> ~/.codex/config.toml</p><p></p><p>Palitan ang dating 1M configuration:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model = "gpt-5.6-sol"</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_context_window = 1000000</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_auto_compact_token_limit = 900000</code></pre><p> </p><p>gamit ang:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model = "gpt-5.6-sol"</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_context_window = 272000</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_auto_compact_token_limit = 250000</code></pre><p> </p><p>Pagkatapos i-save, i-restart ang Codex at gumawa ng bagong session.</p><p></p><p>Kung gusto mo itong pansamantalang ilapat sa isang session lamang, gamitin ang:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code>codex -m gpt-5.6-sol -c model_context_window=272000 -c model_auto_compact_token_limit=250000</code></pre><p> Inirerekomenda rin ang mga sumusunod:</p><p> - Agarang paganahin ang automatic compaction;</p><p> - Hatiin ang mahahabang gawain sa maraming session;</p><p> - Bawasan ang dami ng tool output na ibinabalik nang sabay-sabay;</p><p> - Regular na ibuod at linisin ang mas lumang bahagi ng pag-uusap.</p><p></p><p>Kung talagang kailangan ang 1M-token context window, maaari mong ipagpatuloy ang paggamit ng dating configuration, ngunit ang bahaging lalampas sa 272K ay sisingilin ayon sa mas mataas na tier ng presyo ng OpenAI.</p><p></p><p>Para sa orihinal na gabay sa configuration, sumangguni sa: <a target="_blank" rel="noopener noreferrer nofollow" href="https://ai-router.dev/blog/post-g-677aa3e043912838-how-to-enable-a-1m-token-context-window-in-codex-for-gpt-5-6-sol">Paano paganahin ang 1 milyong Token context window para sa GPT-5.6 Sol sa Codex </a></p><p></p><p>Salamat sa inyong pang-unawa at suporta.</p>
  • 0 Votes
    1 Posts
    203 Views
    A
    <p>According to OpenAI's latest pricing rules, models such as GPT-5.6 Sol use tiered pricing for ultra-long-context requests.</p><p>When the context of a single request exceeds 272K tokens, input, cached reads, and output may be charged according to higher-tier prices. This rule comes from OpenAI's official pricing and is not a temporary markup or abnormal charge from AI-ROUTER.</p><p></p><p>Using GPT-5.6 Sol's standard pricing as an example:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Item Within 272K Over 272K</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ━━━━━━━━━━ ━━━━━━━━━━━━━━━━━━━ ━━━━━━━━━━━━━━━━━</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Input $5 / 1M tokens $10 / 1M tokens</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ────────── ─────────────────── ─────────────────</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Cached read $0.50 / 1M tokens $1 / 1M tokens</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> ────────── ─────────────────── ─────────────────</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> Output $30 / 1M tokens $45 / 1M tokens</code></pre><p></p><p>The final cost is still calculated based on the specific group, account multiplier, and other billing settings.</p><p></p><p>An x2 flag in usage records indicates that the request triggered tiered pricing for long contexts. This flag primarily indicates that input and cached reads entered a higher tier; it does not mean that all charges for the request are simply doubled. Output may use a different tier multiplier.</p><p></p><p><strong> ## How to Avoid Triggering Tiered Pricing</strong></p><p>If you do not need a 1M-token context window, we recommend adjusting Codex's context limit back to 272K or below.</p><p></p><p>Open:</p><p> ~/.codex/config.toml</p><p></p><p>Change the 1M configuration from the original article:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model = "gpt-5.6-sol"</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_context_window = 1000000</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_auto_compact_token_limit = 900000</code></pre><p> </p><p>to:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model = "gpt-5.6-sol"</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_context_window = 272000</code></pre><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code> model_auto_compact_token_limit = 250000</code></pre><p> </p><p>Save the changes, restart Codex, and start a new session.</p><p></p><p>If you only want the change to apply temporarily to a single session, use:</p><pre class="rounded-lg bg-slate-950 px-4 py-3 text-slate-100"><code>codex -m gpt-5.6-sol -c model_context_window=272000 -c model_auto_compact_token_limit=250000</code></pre><p> We also recommend:</p><p> - Enable automatic compaction promptly;</p><p> - Split ultra-long tasks across multiple sessions;</p><p> - Reduce the amount of tool output returned at once;</p><p> - Regularly summarize and clean up older conversation content.</p><p></p><p>If you genuinely need a 1M-token context window, you can continue using the original configuration, but usage beyond 272K will be charged according to OpenAI's higher-tier pricing.</p><p></p><p>For the original configuration instructions, see: <a target="_blank" rel="noopener noreferrer nofollow" href="https://ai-router.dev/blog/post-g-677aa3e043912838-how-to-enable-a-1m-token-context-window-in-codex-for-gpt-5-6-sol">How to Enable a 1M-Token Context Window in Codex for GPT-5.6 Sol </a></p><p></p><p>Thank you for your understanding and support.</p>